Annals of the New York Academy of Sciences
○ Wiley
All preprints, ranked by how well they match Annals of the New York Academy of Sciences's content profile, based on 17 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit. Older preprints may already have been published elsewhere.
Zamorano, A. M.; Haumann, N. T.; Ligato, D.; Maublanc, P.; Brattico, E.; Vuust, P.; Kleber, B.
Show abstract
Musical expertise is often associated with heightened perceptual sensitivity to external sensory stimuli, yet its relationship with internal bodily awareness (interoception) remains elusive. This study examined whether interoceptive ability relates differentially to varying levels of singing expertise and explored if interoception could predispose individuals to musical skills. Professional singers, amateur singers, and non-singers completed a heartbeat discrimination task (interoceptive accuracy; IAcc), self-reported interoceptive sensibility assessment, and comprehensive musical competence measures. Results demonstrated a significant positive association between singing expertise and IAcc, which notably emerged only in professional singers, who significantly outperformed both amateur singers and non-singers. Regression analyses indicated a moderate predictive role of accumulated singing practice for IAcc among trained singers, though emotional awareness partially mediated the relationship between expertise and IAcc. Critically, higher IAcc in non-singers significantly correlated with superior singing accuracy, suggesting enhanced interoception may facilitate singing competence independently of formal vocal training. Altogether, these findings highlight a dual role of interoception, linking it to expert-level singing--partially mediated by emotional awareness-- and independently to musical competence in non-singers. These results reconcile prior inconsistencies by highlighting the embodied nature of singing and the distinct role of musical expertise, while underscoring inherent musically relevant mechanisms beyond practice alone.
Bertolo, M.; Weiss, M. W.; Peretz, I.; Sakata, J.
Show abstract
Isochrony - as in the regular beat of a metronome - is cross-culturally ubiquitous in music. Is this ubiquity due to a widespread biological inclination for acoustic communication having isochronous structure? If so, it should be present in lesser-studied vocal music, in the absence of musical training, and in comparable non-human species vocalizations. We quantified isochrony in an untrained expression of musicality: improvised songs from non-musician adults and children. We also tested for isochrony in songs from zebra finches, a bird that learns complex songs. We analyzed improvised songs from children (n = 38, 3-10 years old) and adult non-musicians (n = 15, 24-82 years old), and songs from juvenile and adult zebra finches (n= 77, [~]50 and 120 days post-hatch, respectively). Isochrony was expressed in non-musicians improvised songs, and in childrens improvised songs to a comparable degree. In contrast, both juvenile and adult zebra finches produced songs with less isochrony than chance. And although zebra finches learn sequences and durations of their songs elements, we found no evidence for the learning of isochrony. These data show that spontaneous isochrony in learned vocalizations appears differently across these two species. We propose that species variation in chorusing behaviors may explain why.
Zioga, I.; Guichet, C.; Di Bernardi Luft, C.; Papageorgiou, C.; Anagnostopoulou, C.
Show abstract
Music is a fundamental medium for human communication, yet it remains unclear whether it enables genuine alignment of internally felt emotions beyond mere emotion perception. Creative expression may be central to this process: when performers improvise with creative intent, they engage in self-expression that could draw listeners into closer emotional and physiological resonance. Here, dual-electrocardiography was recorded across 37 performer-listener dyads while performers generated live improvisations under conventional, unconventional, or creative instructions. We show that creative improvisation enhances dyadic emotional alignment for both valence and arousal, and increases cardiac synchronization. Performer-listener alignment was strongest when performers felt and expressed emotions were themselves aligned, and when listeners showed a gradual arousal buildup. Creativity also elicited higher sublimity, lower unease and vitality, and greater diversity of selected emotional labels. These findings establish creativity as a mechanism through which music achieves one of its most celebrated functions: the creation of shared human experience.
Konvalinka, I.; Sebanz, N.; Knoblich, G.
Show abstract
People synchronize their behavioural and physiological rhythms with each other during social interaction. While interpersonal synchronization has largely been associated with positive effects such as social bonding, some evidence suggests that it may also impair self-regulation and disrupt intrapersonal coordination. Because respiration and heart rhythms are weakly coupled within individuals, we investigated whether synchronizing breathing with another person alters intra-personal cardiorespiratory coupling. Across two experiments, participants synchronized their breathing reciprocally (bidirectional interaction), unidirectionally with a confederate, or with pre-recorded breathing signals, while respiration and electrocardiography (ECG) were continuously measured. Relative phase analyses revealed that bidirectional breathing synchronization induced in-phase synchronization of heart rhythms between individuals. Critically, interpersonal synchronization coincided with cardiorespiratory decoupling: respiration and heart rhythms became more out-of-phase during interaction compared to resting baselines and the unidirectional condition. Moreover, stronger interpersonal respiratory synchronization predicted greater intra-personal cardiorespiratory decoupling, particularly when participants adapted their own breathing to another persons. These findings provide evidence that interpersonal physiological synchronization entails a trade-off with intra-personal physiological coupling, perturbing the phase relationship within ones own physiological system. We propose that aligning ones physiological rhythms with others strengthens self-other coupling but weakens intrapersonal coupling, pointing to a physiological mechanism of self-decoupling during social interaction. Graphical abstractWhen two people synchronize their breathing together, their heart rhythms also synchronize, while their intrapersonal cardiorespiratory rhythms become decoupled, with a perturbed phase-relationship between respiration and cardiac rhythms. Across two experiments, stronger interpersonal respiratory synchronization was associated with more pronounced intrapersonal cardiorespiratory decoupling. These findings suggest a physiological trade-off: aligning bodily rhythms with others strengthens self-other coupling, but may coincide with a phase shift in ones intrinsic cardiorespiratory rhythm. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=96 SRC="FIGDIR/small/679159v2_ufig1.gif" ALT="Figure 1"> View larger version (16K): org.highwire.dtl.DTLVardef@17e4675org.highwire.dtl.DTLVardef@143bf94org.highwire.dtl.DTLVardef@d6b841org.highwire.dtl.DTLVardef@dab5e2_HPS_FORMAT_FIGEXP M_FIG C_FIG
Luck, G.
Show abstract
Degradation of motor control across the adult lifespan due to neurobiological decay is well-established. Correspondences between the dynamics of motor behaviour and the timing of musical performance are also well-documented. In light of the former, the conspicuous absence of age as a mediating factor in investigation of the latter reveals a remarkable gap in our understanding of creative performance across the life course. To examine effects of ageing on musical timing, physical tempo of almost 2000 songs released by top-tier recording artists over their decades-long careers were annotated via a listening and tapping task. A series of regression analyses revealed i) an age-driven downward trend in performance tempo for all artists, ii) significant between-artist variation across time, and iii) within-artist variation that was independent of broader musical trends. Overall, tempo decreased by almost one and a half standard deviations from artists early twenties to their late fifties, a rate of decline comparable to that observed in studies of spontaneous motor tempo. Results are consistent with the slowing-with-age hypothesis, and reveal that, not only is such tempo decline discernible in commercial recordings, the impact of age on tempo is overwhelming for artists most physically connected with their music.
Serre, H.; Harrigian, K.; Park, S.-W.; Sternad, D.
Show abstract
Rhythmic ability has been studied for more than a century in laboratory settings testing timed finger taps. While robust results emerged, it remains unclear whether these findings reflect behavioral limitations in realistic scenarios. This study tested the synchronization-continuation task in a museum with 455 visitors of a wide variety of ages (5-74yrs), musical experiences (0-40yrs) and educational and cultural backgrounds. Adopting a dynamic systems perspective, three metronome pacing periods were anchored around each individuals preferred tempo, and 20% faster and 20% slower. Key laboratory findings were replicated and extended: timing error and variability decreased during childhood and increased in older adults and were lower, even with moderate musical experience. Consistent with an oscillator perspective, timing at non-preferred tempi drifted toward their preferred rate. Overall, these findings demonstrate that timing limitations may reflect attractor properties of a neural oscillator and its signature is still present even in noisy, naturalistic settings.
MAKRIS, S.; Langley, B.; Page, R.; Perris, E.; Karkou, V.; Cazzato, V.
Show abstract
BackgroundMirroring is a foundational method used in Dance Movement Therapy (DMT), assumed to foster empathy and therapeutic attunement, yet its embodied dynamics remain insufficiently studied. In this paper, we provide the first quantitative exploration of client-therapist mirroring across dyadic and triadic formats, examining how synchrony unfolded during a structured mirroring exercise in which participants alternated between leading and following roles. MethodologyUsing optical motion capture and time-series modelling, we quantified movement coordination in dyadic (female client-therapist; male client-therapist) and triadic (therapist with both clients) interactions. ResultsIn dyadic tasks, the female client-therapist interaction was marked by tight temporal alignment, significant synchrony, robust predictive accuracy, and clear client-to-therapist influence, consistent with kinaesthetic empathy and affect-sensitive entrainment. By contrast, the male client-therapist dyad exhibited weaker and more delayed temporal coupling, alongside reduced phase synchronisation and fewer directional dependencies, despite comparable levels of interpersonal proximity. In the triadic task, temporal entrainment attenuated: therapist movement had few matching qualities to clients movement, yet recurrent synchrony with both clients persisted, suggesting a strategic shift from fine-grained entrainment to stable postural scaffolding under divided attention. DiscussionThese findings demonstrated that mirroring is not a uniform technique, but a family of embodied coordination modes flexibly recruited according to relational context and client expressivity. They align with theories of embodied simulation and affect attunement, implicating rapid motor resonance in dyadic entrainment and interoceptive-affective scaffolding in triadic stability. Clinically, the results underscore the need for training in flexible embodied strategies, split attention, and equitable allocation of attunement in group work. More broadly, they open a translational agenda linking kinematic synchrony to neural, interoceptive, and autonomic mechanisms, positioning mirroring as both an experiential hallmark and a measurable mechanism of change in embodied psychotherapy.
Campo, F. F.; Carlomagno, F.; Yilmaz, B.; Quaranta, L.; Carraturo, G.; Rivolta, D.; Brattico, E.
Show abstract
While sounds are essential for human development, existing research primarily emphasizes visual and spatial memory and speech in relation to auditory processes, with a lack of a cohesive understanding of memory mechanisms for non-verbal sounds. This systematic review and coordinate-based meta-analysis of neuroimaging findings aims to comprehensively organize the literature for identifying the key brain mechanisms involved in short-term memory, working memory, and long-term memory for non-verbal auditory information. Additionally, we aimed to identify whether and how individual differences in neural memory processes related to auditory expertise (musicianship), aging or auditory impairments such as amusia are explored in the literature. Our review included ninety studies meeting the selection criteria, with only thirteen studies containing brain coordinates could be included in the meta-analysis. The coordinate-based meta-analysis identified a frontal hub for non-verbal auditory memory encompassing the medial frontal gyrus, cingulate gyrus and superior frontal gyrus.
Proverbio, A. M.; Qin, C.
Show abstract
This study examines the temporal dynamics of expressive piano performance by means of a quantitative analysis of motor timing in an elite pianist, with particular reference to stylistic contrasts between Baroque and Romantic repertoire. In line with kinematic models of expressive timing, which describe musical performance as reflecting principles of biological motion, we examined whether a common temporal structure underlies stylistically divergent executions. Despite marked differences in structural complexity and gesture density, both performances exhibited a shared low-frequency oscillatory pattern ([~]0.36 Hz) in beat-level timing variability. This infra-delta rhythmic modulation is consistent with the presence of an underlying motor timing scaffold and suggests a common temporal organization across expressive behaviors. These findings support the hypothesis that musical performance relies on a rhythmically structured control architecture, potentially shared with other complex motor activities such as speech and locomotion.
Shembel, A.; Young, E.; Smeltzer, J.; Behroozmand, R.
Show abstract
PurposeAuditory and somatosensory systems jointly support vocal motor control, yet their relative and combined contributions in primary muscle tension dysphonia (pMTD) remain poorly understood. This study examined how auditory and somatosensory feedback influence vocal motor control and adaptation in individuals with and without pMTD. MethodFifty-one participants (pMTD: n = 10; controls: n = 41) completed a 200-trial altered auditory feedback (AAF) paradigm involving sustained vowel productions with fundamental frequency (f) downward shifts at -100 cents. The task was repeated with nebulized lidocaine to transiently attenuate laryngeal somatosensory input. Data were analyzed to extract the magnitude and direction of vocal adjustments across baseline, ramp, hold, and washout phases. ResultsVocally healthy controls demonstrated robust adaptive vocal response to f feedback alterations followed by an adaptive after-effect during the washout phase, particularly when somatosensation was intact. In contrast, when somatosensation was intact, individuals with pMTD displayed exaggerated overshooting vocal responses in f, lacked typical after-effect adaptation, and showed greater variability overall. Notably, somatosensory disruption stabilized adaptation in pMTD with nebulized lidocaine. ConclusionThese findings demonstrate that pMTD is characterized by maladaptive reliance on somatosensory input that interferes with normal auditory-motor integration. While vocally healthy speakers use somatosensory cues to stabilize vocal control, in pMTD these cues may act as a source of instability. The results underscore the need for therapeutic approaches targeting sensory-motor integration, including individualized interventions that recalibrate the balance between auditory and somatosensory feedback. This study reframes pMTD as a multisensory integration disorder and opens new directions for mechanism-driven voice therapy.
McPherson-McNato, M.; Undurraga, E.; Seidle, A.; Honeycutt, O.; McDermott, J. H.
Show abstract
Pitch is a building block of speech and music, but the extent to which pitch perception is shared across cultures is unclear. Evidence from Western participants suggests that pitch perception relies on multiple representations. For instance, harmonic tones are easier to discriminate in noise than inharmonic tones despite comparable discrimination in quiet, suggesting that different representations are used in noise and quiet. We tested whether these effects are present cross-culturally, comparing participants from the US and a Bolivian Amazonian Indigenous community (Tsimane). Participants heard two-note melodies and reproduced the melody by singing. Tones were either harmonic or inharmonic and were presented in noise or quiet. Both groups exhibited two characteristics of pitch perception previously seen in US listeners: the direction of pitch changes could be reproduced with equal accuracy for harmonic and inharmonic tones in quiet but was better for harmonic than inharmonic tones in noise. However, replicating previous work, Tsimane vocal reproductions were much less likely to be related to the absolute pitch or chroma of the stimulus notes, differing from the tendency seen in Western participants to match pitch and/or chroma. Pitch and chroma matching behavior were more prominent in a subset of Tsimane whose responses to a demographic survey suggested greater integration with global and Bolivian markets and culture. The results demonstrate that the basic structure of pitch perception is shared across cultures despite other differences in pitch-related behavior that are plausibly driven by culture-specific experience.
Barumerli, R.; Geronazzo, M.; Cesari, P.
Show abstract
Our brain maps the space immediately surrounding the body, the peripersonal space (PPS), to sharpen sensory-motor coordination whenever an object enters it. Within PPS, past research demonstrated how several factors influence motor readiness: from exogenous factors, such as body-object distance and stimulus semantics, to endogenous traits like personality traits. Nevertheless, most paradigms rely on vision or touch, relegating hearing to a supporting role and leaving auditory-only contributions unclear. Here, we tested whether affective content and individual traits modulate motor planning for looming sounds that stop within PPS. Thirty-three adults completed three auditory-only tasks in which positive, negative, or neutral sounds halted at five simulated distances from the participants ears (0.3-0.7 m). We recorded anticipatory postural adjustments, distance estimates, affective ratings, and sensory suggestibility via a questionnaire. Motor responses were largely anticipated as sounds stopped nearer the body, while delayed and less precise for semantic (positive or negative) than neutral sounds. Higher suggestibility predicted longer and more variable premotor latencies, particularly for non-semantic sounds. These findings show that auditory cues alone engage flexible sensorimotor mechanisms within PPS, where exogenous (distance, semantics) and endogenous (suggestibility) factors jointly shape motor readiness and spatial perception.
Jachs, B.; Garcia, M. C.; Canales-Johnson, A.; Bekinschtein, T.
Show abstract
Subjective experiences are hard to capture quantitatively without losing depth and nuance, and subjective report analyses are time-consuming, their interpretation contested. We describe Temporal Experience Tracing, a method that captures relevant aspects of the unified conscious experience over a continuous period of time. The continuous multidimensional description of an experience allows us to computationally reconstruct common experience states. Applied to data from 852 meditations - from novice (n=20) and an experienced (n=12) meditators practising Breathing, Loving-Kindness and Open-Monitoring meditation - we reconstructed four recurring experience states with an average duration of 6:46 min (SD = 5:50 min) and their transition dynamics. Three of the experience states assimilated the three meditation styles practiced, and a fourth experience state represented a common low-motivational, off-task state for both groups. We found that participants in both groups spent more time in the task-related experience state during Loving Kindness meditation than other meditation styles and were less likely to transition into an off-task experience state during Loving Kindness meditation than during Breathing meditation. We demonstrate that drawing the dynamics of experience enables the quantitative analysis of subjective experiences, transforming the time dimension of the stream of consciousness from narrative to measurable.
Matsui, T.; Yamaguchi, T.; Funabashi, D.; Takahashi, S.; Matsuoka, H.; Dobashi, S.; Kigoshi, K.; Yamada, S.; Takagi, H.
Show abstract
Oxytocin has been proposed as a pharmacological target for social dysfunction. However, clinical trials of exogenous oxytocin have yielded modest and inconsistent benefits, highlighting the need for naturalistic, non-pharmacological approaches that recruit endogenous oxytocin. This prospective observational field study examined whether live sports spectating in a modest collegiate setting engages endogenous oxytocin and heart rate synchrony, and how these processes relate to social bonding. Sixty casual spectators at two university basketball games provided salivary samples for oxytocin and cortisol and completed psychological questionnaires before the game, at halftime, and after the game, while heart rate was continuously recorded. Oxytocin dynamics were baseline dependent: individuals with low pre-game oxytocin showed clear increases during spectating, whereas those with high pre-game oxytocin maintained elevated concentrations; across both groups, cortisol decreased and the oxytocin-to-cortisol ratio increased, indicating a shift toward an oxytocin-dominant, low-stress hormonal profile. Live spectating also enhanced heart rate synchrony between spectators, and higher oxytocin levels and stronger synchrony were associated with greater enjoyment, stronger unity with players and fellow fans, higher flow, and stronger intention to revisit. Mediation analyses indicated that perceived unity mediated links between physiological responses and experiential outcomes. These findings identify coordinated oxytocin dynamics and physiological synchrony as candidate mechanisms through which live sports spectating can foster social connection and positive psychological states in everyday, low-stakes settings. Live sports events may represent a simple, scalable non-pharmacological approach to engaging oxytocin-related systems and interpersonal synchrony to promote social connection in everyday life.
Whiteford, K. L.; Baltzell, L. S.; Chiu, M.; Cooper, J. K.; Faucher, S.; Goh, P. Y.; Hagedorn, A.; Irsik, V. C.; Irvine, A.; Lim, S.-J.; Mesik, J. L.; Mesquita, B.; Oakes, B.; Rajappa, N.; Roverud, E.; Schrlau, A. E.; Van Hedger, S. C.; Bharadwaj, H. M.; Johnsrude, I. S.; Kidd, G.; Luebke, A. E.; Maddox, R. K.; Marvin, E. W.; Perrachione, T. K.; Shinn-Cunningham, B. G.; Oxenham, A. J.
Show abstract
Musical training has been associated with enhanced neural processing of sounds, as measured via the frequency following response (FFR), implying the potential for human subcortical neural plasticity. We conducted a large-scale multi-site preregistered study (n > 260) to replicate and extend the findings underpinning this important relationship. We failed to replicate any of the major findings published previously in smaller studies. Musical training was related neither to enhanced spectral encoding strength of a speech stimulus (/da/) in babble nor to a stronger neural-stimulus correlation. Similarly, the strength of neural tracking of a speech sound with a time-varying pitch was not related to either years of musical training or age of onset of musical training. Our findings provide no evidence for plasticity of early auditory responses based on musical training and exposure.
Oku, T.; Makimoto, Y.; Shioki, M.; Koike, H.; Furuya, S.
Show abstract
Remote instruction is increasingly used to teach complex sensorimotor skills, yet conventional audio-video communication poorly conveys the fine-grained attentional cues that support expert guidance. This study tested whether real-time bidirectional gaze sharing enhances remote transfer of piano performance skill by restoring joint visual attention between teacher and learner. Twenty-seven conservatory-level pianists were randomly assigned either to a group, in which teacher and learner gaze positions were visualized during online instruction, or to a group receiving otherwise identical instruction without gaze cues. We recorded eye movements with wearable eye trackers and evaluated piano performance using a high-resolution key-motion sensing system. Real-time gaze sharing increased learners gaze-pattern similarity to a teacher, which was not evident in the control group. A parallel effect was observed for head-movement similarity. Critically, gaze sharing also reduced variability of the key-descending velocity at the moment of finger-key contact for the right-hand landing after a leap, a feature associated with unstable key-striking velocity. These findings exhibit that gaze information is not merely an auxiliary communication cue but a timing-critical coordination channel for remote motor instruction. By augmenting video-mediated pedagogy with shared attentional dynamics, the proposed system offers a framework for transmitting tacit, high-dexterity skills across distance.
Galvez-Pol, A.; Rambaud, V.; Christensen, J. F.; Kilner, J. M.
Show abstract
In non-verbal communication, observers infer emotions from visible facial movements, yet emotional experiences are described in internal bodily terms (e.g., "my heart skipped a beat"). This contrast highlights a tension between external sensory cues and internal signals. In this context, we examined an overlooked gap in affective science: what makes an emotional portrayal believable, and do believability judgments reflect only what observers can see or also the portraying person's internal cardiac dynamics? To test this, we created 311 scenario-driven acting clips designed to avoid prototypical posed displays. For each clip, we quantified facial movement magnitude from the video, recorded ECG during preparation and enactment, and collected actors' self-reports. Online participants (N = 371) viewed these clips and provided emotion recognition responses and continuous ratings of believability, valence, or arousal. The results show that believability decreased as movement magnitude increased, with a non-linear relationship indicating a stronger penalty as motion increased. Valence further shaped this pattern, with increasing movement reducing believability more strongly for portrayals with negative valence. This effect persisted after accounting for intended emotion, perceived arousal, and emotion recognizability. Cardiac dynamics varied during performance, and actors' higher heart rate variability was associated with higher believability for positively valenced portrayals. Together, these findings show that believability is driven by visible movement cues interpreted in relation to valence, with actors' cardiac dynamics showing selective alignment with believability. These results identify core components of believable emotional expressions and provide a basis for studying such judgments in everyday social interaction.
ADL ZARRABI, A.; Jeulin, M.; Bardet, P.; Commere, P.; Naccache, L.; Aucouturier, J.-J.; Ponsot, E.; Villain, M.
Show abstract
After a right hemisphere stroke, more than half of the patients are impaired in their capacity to produce or comprehend speech prosody. Yet, and despite its social-cognitive consequences for patients, aprosodia following stroke has received scant attention. In this report, we introduce a novel, simple psychophysical procedure which, by combining systematic digital manipulations of speech stimuli and reverse-correlation analysis, allows estimating the internal sensory representations that subtend how individual patients perceive speech prosody, and the level of internal noise that govern behavioral variability in how patients apply these representations. Tested on a sample of N=22 right-hemisphere stroke survivors and N=12 healthy controls, the representation+noise model provides a promising alternative to the clinical gold standard for evaluating aprosodia (MEC): both parameters strongly associate with receptive, and not expressive, aprosodia measured by MEC within the patient group; they have better sensitivity than MEC for separating high-functioning patients from controls; and have good specificity with respect to non-prosody-related impairments of auditory attention and processing. Taken together, individual differences in either internal representation, internal noise, or both, paint a potent portrait of the variety of sensory/cognitive mechanisms that can explain impairments of prosody processing after stroke.
Segura, E.; Lorenzo-Seva, U.; Zatorre, R.; Kleber, B. A.; Rodriguez-Fornells, A.
Show abstract
Singing is an innate human behaviour present across cultures and the lifespan. Despite lacking direct biological advantages, its ubiquity suggests that it is intrinsically rewarding. This research aimed to investigate the underlying factors that explain variability in sensitivity to deriving reward and enjoyment from natural singing in the general population. In Study 1 (n = 606), an initial pool of items describing daily, non-professional singing behaviours were administered to an international adult sample. Exploratory factor analysis revealed a unidimensional structure of 20 items with acceptable model fit, organized into five facets representing distinct domains of singing-related rewards: 1) pleasure and emotional evocation, 2) social singing reward, 3) singing frequency, 4) mood regulation through singing, and 5) inattentional singing during routine tasks. In Study 2 (n = 430), confirmatory factor analysis in a new sample supported this structure. When both samples were combined (n = 1036), the unidimensional model defined by these five facets showed acceptable to excellent goodness-of-fit indices, supporting the conceptualization of singing reward as a multidimensional construct with differentiated facets. This led to the Barcelona-Aarhus Natural Singing Engagement Questionnaire (BANSEQ), which demonstrated excellent reliability ( = .94) and population-level stability. Study 3 (n = 1036) tested the convergent validity of BANSEQ with measures of music reward and engagement and identified sociodemographic and psychological correlates across the five facets of singing reward. Overall, these findings characterize the sources of individual differences in the hedonic experience of natural singing and propose BANSEQ as a robust psychometric tool for its assessment in the general population.
Barbieri, P.; Sarasso, P.; Rossi Sebastiano, A.; Frascaroli, J.; Poles, K.; Peila, C.; Coscia, A.; Garbarini, F.; Ronga, I.
Show abstract
Isolating relevant sounds in the auditory stream is a crucial feature accomplished by human infants and a pivotal ability for language acquisition. Therefore, it is reasonable to postulate the existence of early mechanisms reorienting attention toward salient acoustic stimuli. Previous studies suggest that infants consider consonant sounds as more salient than dissonant ones, because the former resemble human vocalizations. However, systematic evidence investigating the neural processes underlying consonance tuning in newborns is still scarce. Here, we investigate newborns ability to recognize and learn salient auditory stimuli by collecting Mismatch Responses (MMRs) to consonant and dissonant sounds and by computing the trial-by-trial correlation of the neural signal with Bayesian Surprise (a theoretical measure of learning). We present 22 healthy newborns (40.4 {+/-} 15.8 hours) with a pseudo-random sequence of deviant and standard auditory events, while we record their electroencephalogram. Our results show that newborns exhibit a neural encoding of auditory regularities for all sound types (consonant and dissonant), as demonstrated by the presence of MMRs and significant correlation of the neural signal with Bayesian Surprise. Furthermore, consonant and dissonant sounds elicited MMRs and correlations with Bayesian Surprise of opposite polarities, with consonant auditory stimulation evoking negative responses, reminiscent of an adult-like MMR. Overall, our findings suggest that newborns display a dedicated perceptual learning mechanism for salient consonant sounds. We speculate that this mechanism might represent an evolutionary-achieved neural tuning to detect and learn salient auditory stimuli with acoustic features resembling human vocalizations. SIGNIFICANCE STATEMENTDiscriminating salient sounds in noisy sensory streams is a fundamental ability displayed by human infants, pivotal for acquiring crucial skills including language. Our study shed light on this ability by: (1) investigating perceptual learning mechanisms in newborns with a neurocomputational approach; (2) exploring the role of salient consonant sounds in modulating such mechanisms. Since human vocalizations are often consonant, the presence of a mechanism dedicated to enhance the processing of consonant sounds in newborns would confer evolutionary advantages. Our findings, indicating that newborns possess a dedicated and more refined perceptual learning mechanism to process consonance, corroborates this hypothesis. We speculate that this neural mechanism might facilitate the identification of salient acoustic input and support language acquisition in early infancy.